Alibaba Qwen Launches Qwen-Audio-3.1: Five Speech Models Released Simultaneously, ASR Prices Drop by 95%
Qwen launches the Qwen-Audio-3.1 series of large speech models, with comprehensive upgrades in speech recognition, synthesis, and real-time interaction, and introduces the audio creation model TTS-Next and the understanding model ASR-Next. Five models cover the complete capability stack of 'understanding - generation - interaction - creation'. At the same time, all models are discounted, with TTS dropping by about 70% to lower the usage threshold.